OpenAI Research
Aug 07, 2026
Responding to the next frontier of critical cyber capabilities
OpenAI says preliminary internal and expert assessments of its upcoming Astra model mean it cannot rule out the Critical cyber-capability threshold in its Preparedness Framework. The company says it has paused certain internal work pending stronger controls, including isolated testing, restricted network and tool access, enhanced weight protection, monitoring, and external testing; these are provider statements, not an independent capability verification.
- OpenAI says it cannot rule out that its upcoming Astra model has reached the Critical cybersecurity threshold in its Preparedness Framework.
- OpenAI says it is pausing internal Astra activities that do not meet strengthened security-control requirements while it expands testing.
Why it mattersA frontier developer publicly treating a model as potentially critical for cyber risk is a material governance and deployment signal. It raises the importance of independent evaluation and of safeguards that can be audited before broad agentic deployment.
Anthropic Research
Aug 07, 2026
Improving Fable 5's biology safeguards
Anthropic says it updated Claude Fable 5’s biology-safety classifier to reduce benign-query fallbacks by about 85% in its testing, while continuing to route dual-use professional biology and drug-development requests to a less capable model. The reported reduction is a provider evaluation; the continuing restrictions and trusted-access plans show that access is still deliberately constrained.
- Anthropic reports that its classifier update reduced biology-related Fable 5 fallbacks by about 85% across product surfaces in its testing.
- Anthropic says dual-use professional biology and drug-development requests will continue to be routed away from Fable 5 pending trusted-access pathways.
Why it mattersThe update illustrates the operational trade-off in frontier-model deployment: lowering false positives can expand useful scientific and clinical assistance, but only if misuse controls remain robust. It is a product-access and safety-governance signal, not independent evidence that the safeguards are sufficient.